Center-Based Indexing in Vector and Metric Spaces

نویسنده

  • Arkadiusz Wojna
چکیده

The paper addresses the problem of indexing data for k nearest neighbors (k-nn) search. Given a collection of data objects and a similarity measure the searching goal is to find quickly the k most similar objects to a given query object. We present a top-down indexing method that employs a widely used scheme of indexing algorithms. It starts with the whole set of objects at the root of an indexing tree and iteratively splits data at each level of indexing hierarchy. In the paper two different data models are considered. In the first, objects are represented by vectors from a multi-dimensional vector space. The second, more general, is based on an assumption that objects satisfy only the axioms of a metric space. We propose an iterative k-means algorithm for tree node splitting in case of a vector space and an iterative k-approximate-centers algorithm in case when only a metric space is provided. The experiments show that the iterative k-means splitting procedure accelerates significantly k-nn searching over the one-step procedure used in other indexing structures such as GNAT, SS-tree and M-tree and that the relevant representation of a tree node is an important issue for the performance of the search process. We also combine different search pruning criteria used in BST, GHT nad GNAT structures into one and show that such a combination outperforms significantly each single pruning criterion. The experiments are performed for benchmark data sets of the size up to several hundreds of thousands of objects. The indexing tree with the k-means splitting procedure and the combined search criteria is particularly effective for the largest tested data sets for which this tree accelerates searching up to several thousands times.

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

منابع مشابه

Some properties of continuous linear operators in topological vector PN-spaces

The notion of a probabilistic metric  space  corresponds to thesituations when we do not know exactly the  distance.  Probabilistic Metric space was  introduced by Karl Menger. Alsina, Schweizer and Sklar gave a general definition of  probabilistic normed space based on the definition of Menger [1]. In this note we study the PN spaces which are  topological vector spaces and the open mapping an...

متن کامل

Concurrent vector fields on Finsler spaces

In this paper, we prove that a non-Riemannian isotropic Berwald metric or a non-Riemannian (α,β) -metric admits no concurrent vector fields. We also prove that an L-reducible Finsler metric admitting a concurrent vector field reduces to a Landsberg metric.In this paper, we prove that a non-Riemannian isotropic Berwald metric or a non-Riemannian (α,β) -metric admits no concurrent vector fi...

متن کامل

Common fixed point results on vector metric spaces

In this paper we consider the so called a vector metric space, which is a generalization of metric space, where the metric is Riesz space valued. We prove some common fixed point theorems for three mappings in this space. Obtained results extend and generalize well-known comparable results in the literature.

متن کامل

System of fuzzy fractional differential equations in generalized metric space

In this paper, we study the existence of integral solutions of fuzzy fractional differential systems with nonlocal conditions under Caputo generalized Hukuhara derivatives. These models are considered in the framework of completegeneralized metric spaces in the sense of Perov. The novel feature of our approach is the combination of the convergentmatrix technique with Schauder fixed point princi...

متن کامل

Some Fixed Point Theorems in Generalized Metric Spaces Endowed with Vector-valued Metrics and Application in Linear and Nonlinear Matrix Equations

Let $mathcal{X}$ be a partially ordered set and $d$ be a generalized metric on $mathcal{X}$. We obtain some results in coupled and coupled coincidence of $g$-monotone functions on $mathcal{X}$, where $g$ is a function from $mathcal{X}$ into itself. Moreover, we show that a nonexpansive mapping on a partially ordered Hilbert space has a fixed point lying in  the unit ball of  the Hilbert space. ...

متن کامل

Vector Valued multiple of $chi^{2}$ over $p$-metric sequence spaces defined by Musielak

In this article, we define the vector valued multiple of $chi^{2}$ over $p$-metric sequence spaces defined by Musielak and study some of their topological properties and some inclusion results.

متن کامل

ذخیره در منابع من


  با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید

برای دانلود متن کامل این مقاله و بیش از 32 میلیون مقاله دیگر ابتدا ثبت نام کنید

ثبت نام

اگر عضو سایت هستید لطفا وارد حساب کاربری خود شوید

عنوان ژورنال:
  • Fundam. Inform.

دوره 56  شماره 

صفحات  -

تاریخ انتشار 2003